Goto

Collaborating Authors

 deploying ml solution


Deploying ML solutions with low latency in Python

#artificialintelligence

When we aim for better accuracies, sometimes we forget that the algorithms become more massive and slower. How do you deploy your solution? Which framework to use? Can you use Python for deploying my solution? If you are curious to solve these questions, join me in this talk to discover TensorRT and DeepStream and how they reduce your algorithm's latency and memory footprint. NVIDIA TensorRT is an SDK for high-performance deep learning inference.